Using Genetic Algorithms with Lexical Chains for Automatic Text Summarization
نویسندگان
چکیده
Automatic text summarization takes an input text and extracts the most important content in the text. Determining the importance of information depends on several factors. In this paper, we combine two different approaches that have been used in the text summarization domain. The first one is using genetic algorithms to learn the patterns in the documents that lead to the summaries. The other one is using lexical chains as a representation of the lexical cohesion that exists throughout the text. We propose a novel approach that incorporates lexical chains into the model as a feature and learns the feature weights using genetic algorithms. The experiments performed on the CAST corpus showed that combining different classes of features and also including lexical chains outperform the classical approaches.
منابع مشابه
Computing Lexical Chains for Automatic Arabic Text Summarization
Automatic Text Summarization has received a great deal of attention in the past couple of decades. It has gained a lot of interest especially with the proliferation of the Internet and the new technologies. Arabic as a language still lacks research in the field of Information Retrieval. In this paper, we explore lexical cohesion using lexical chains for an extractive summarization system for Ar...
متن کاملLexical Cohesion Based Topic Modeling for Summarization
In this paper, we attack the problem of forming extracts for text summarization. Forming extracts involves selecting the most representative and significant sentences from the text. Our method takes advantage of the lexical cohesion structure in the text in order to evaluate significance of sentences. Lexical chains have been used in summarization research to analyze the lexical cohesion struct...
متن کاملEfficiently Computed Lexical Chains as an Intermediate Representation for Automatic Text Summarization
While automatic text summarization is an area that has received a great deal of attention in recent research, the problem of efficiency in this task has not been frequently addressed. When considering the size and quantity of documents available on the Internet and from other sources, the need for a highly efficient tool that produces usable summaries is clear. We present a linear time algorith...
متن کاملAutomatic Text Summarization Using Lexical Chains: Algorithms and Experiments
Summarization is a complex task that requires understanding of the document con tent to determine the importance of the text. Lexical cohesion is a method to identify connected portions of the text based on the relations between the words in the text. Lexical cohesive relations can be represented using lexical chains. Lexical chains are sequences of semantically related words spread over the e...
متن کاملAn Automatic Text Summarization Using Lexical Cohesion and Correlation of Sentences
Due to substantial increase in the amount of information on the Internet, it has become extremely difficult to search for relevant documents needed by the users. To solve this problem, Text summarization is used which produces the summary of documents such that the summary contains important content of the document. This paper proposes a better approach for text summarization using lexical chai...
متن کامل